Fix double-tracing in SpecPropPass #15485

GregoryComer · 2025-10-31T00:23:36Z

Our current SpecPropPass doesn't properly capture the effect of guards in the shape environment due to double-tracing certain ops. The problem looks like this:

Every time we trace through the graph, we generate new symints.
That's fine, since shape_env will pick up guards during the retrace.
Problem is that SpecPropPass does this twice. Once to generate the spec and then once by calling super().call_operator(...) (https://github.com/.../exir/passes/spec_prop_pass.py...).
The tensor spec gets the symint from the first. But the graph and guards use the second.
Hence the tensor spec doesn't pick up on guards.

To resolve this, I've updated the SpecPropPass to re-trace the graph and then generate specs based on the meta values, not the traced ProxyValues (thanks @angelayi for the suggestion). This resolves the issue.

I originally saw this issue with the NMS torchvision op, but to avoid adding a new dep to the core EXIR tests, I've written a test with a custom op that uses an unbacked symint in the meta kernel output shape to replicate the bug in the same way.

Differential Revision: D85913581

pytorch-bot · 2025-10-31T00:23:39Z

🔗 Helpful Links

🧪 See artifacts and rendered test results at hud.pytorch.org/pr/pytorch/executorch/15485

📄 Preview Python docs built from this PR

Note: Links to docs will display an error until the docs builds have been completed.

✅ You can merge normally! (1 Unrelated Failure)

As of commit 7da6e25 with merge base 18c1c5b ():

BROKEN TRUNK - The following job failed but were present on the merge base:

👉 Rebase onto the `viable/strict` branch to avoid these failures

pull / android / run-emulator (gh) (trunk failure)
Timeout waiting for emulator to boot.

This comment was automatically generated by Dr. CI and updates every 15 minutes.

meta-codesync · 2025-10-31T00:23:45Z

@GregoryComer has exported this pull request. If you are a Meta employee, you can view the originating Diff in D85913581.

Summary: Pull Request resolved: #15485 Differential Revision: D85913581

Summary: Pull Request resolved: pytorch#15485 Differential Revision: D85913581

GregoryComer · 2025-11-07T00:38:23Z

One additional note is that aliasing analysis feels pretty fragile as is. I fixed several subtle issues (luckily caught by CI) where my changes were accidentally planning two seperate tensors when they should alias / share one TensorSpec.

I'm wondering if we should re-write this pass again to either rely on ProxyValue reference equality or otherwise introduce some proper aliasing analysis. This is as opposed to hard coding that getitem and output, for example, always alias their argument.

This seems like it could get messy with non-functional custom ops or defunctionalization, in general. @JacobSzwejbka @angelayi what are your thoughts on this?

exir/passes/spec_prop_pass.py

Summary: Pull Request resolved: pytorch#15485 Differential Revision: D85913581

meta-codesync · 2025-11-11T00:48:35Z

@GregoryComer has imported this pull request. If you are a Meta employee, you can view this in D85913581.

Summary: Pull Request resolved: pytorch#15485 Differential Revision: D85913581

GregoryComer · 2025-11-13T02:34:54Z

Note that the moshi and zephyr size test failures are pre-existing.

Copilot

Pull request overview

This PR fixes a double-tracing issue in SpecPropPass where tensor specs were generated with symints from a different trace than the one used for guards, causing guards on unbacked symints to be lost. The fix refactors SpecPropPass to perform a single re-trace using the parent ExportPass class and then generate specs from the resulting metadata, ensuring consistency between specs and guards.

Key changes:

Rewrote SpecPropPass.__call__() to re-trace once and populate specs from meta values
Removed individual node handler methods (placeholder, call_operator, call_getitem, etc.) in favor of unified spec generation
Added test case with custom op using unbacked symints to verify guard propagation

Reviewed changes

Copilot reviewed 2 out of 2 changed files in this pull request and generated 3 comments.

File	Description
exir/passes/spec_prop_pass.py	Complete rewrite of SpecPropPass to use single re-trace strategy; replaces per-node callbacks with post-trace spec generation from meta values
exir/tests/test_passes.py	Adds custom ops (unbacked, unbacked.out) and test case to verify spec propagation correctly captures guards for unbacked symints

💡 Add Copilot custom instructions for smarter, more guided reviews. Learn how to get started.

exir/passes/spec_prop_pass.py

exir/tests/test_passes.py

larryliu0820 · 2025-12-04T19:03:06Z

exir/passes/spec_prop_pass.py

+                        if "spec" not in node.meta:
+                            node.meta["spec"] = pytree.tree_map(make_spec, meta_val)
+                    else:
+                        node.meta["spec"] = pytree.tree_map(make_spec, meta_val)


Consider merging these 2 conditions?

There is some weird existing behavior here that seems to need to be preserved (barring a larger update). Basically, we don't want to regenerate call_delegate node specs but do want to regenerate everything else. I'll add a comment detailing why.

larryliu0820 · 2025-12-04T19:04:18Z

exir/passes/spec_prop_pass.py

-    def call_cond(self, pred, true_fn, false_fn, inputs, meta):
-        # true_fn/false_fn return tensors of the same shape, so we can pick
-        # either one here.
-        *_, true_out_node = true_fn.graph.nodes
-        meta["spec"] = pytree.tree_map(make_spec, true_out_node.meta["val"])
-        return super().call_cond(pred, true_fn, false_fn, inputs, meta)
-
-    def call_while(
-        self,
-        cond_fn: torch.fx.GraphModule,
-        body_fn: torch.fx.GraphModule,
-        carried_inputs: List[ProxyValue],
-        additional_inputs: List[ProxyValue],
-        meta: NodeMetadata,
-    ):
-        meta["spec"] = pytree.tree_map(make_spec, carried_inputs)
-        return super().call_while(
-            cond_fn, body_fn, carried_inputs, additional_inputs, meta
-        )


Why we don't have to handle condition and while anymore?

They should be handled by having the tracing logic use ExportPass to regenerate the meta values and then assigning spec values for each node correspondingly. I did go ahead and specific tests for cond and while to verify that the specs are generated correctly. As long as the cond + while outputs don't alias anything else (my understanding is that this should be the case), it should be good.

Summary: Our current SpecPropPass doesn't properly capture the effect of guards in the shape environment due to double-tracing certain ops. The problem looks like this: * Every time we trace through the graph, we generate new symints. * That's fine, since shape_env will pick up guards during the retrace. * Problem is that SpecPropPass does this twice. Once to generate the spec and then once by calling super().call_operator(...) ([https://github.com/.../exir/passes/spec_prop_pass.py...](https://github.com/pytorch/executorch/blob/11f752cf84b296a39c0b74b889d618f279bc8186/exir/passes/spec_prop_pass.py#L98)). * The tensor spec gets the symint from the first. But the graph and guards use the second. * Hence the tensor spec doesn't pick up on guards. To resolve this, I've updated the SpecPropPass to re-trace the graph and then generate specs based on the meta values, not the traced ProxyValues (thanks angelayi for the suggestion). This resolves the issue. I originally saw this issue with the NMS torchvision op, but to avoid adding a new dep to the core EXIR tests, I've written a test with a custom op that uses an unbacked symint in the meta kernel output shape to replicate the bug in the same way. Differential Revision: D85913581 Pulled By: GregoryComer

larryliu0820 · 2025-12-04T22:22:40Z

exir/passes/spec_prop_pass.py

+                    elif (
+                        node.op == "call_function"
+                        and node.target == executorch_call_delegate
+                    ):
+                        # Note: We currently rely on delegate node specs not being regenerated,
+                        # as the spec is set somewhat manually when adding the call delegate node.
+                        # If we regenerate, it can change and break lowering (it becomes a tuple?).
+                        # Ideally, we should figure out how to make the spec regeneration not break
+                        # things.
+                        #
+                        # We do need to regenerate non-call-delegate node specs, as this pass is called
+                        # multiple times in some lowering paths (backends can and do call it).
+                        if "spec" not in node.meta:
+                            node.meta["spec"] = pytree.tree_map(make_spec, meta_val)
+                    else:
+                        node.meta["spec"] = pytree.tree_map(make_spec, meta_val)


I meant something like:

Suggested change

elif (

node.op == "call_function"

and node.target == executorch_call_delegate

):

# Note: We currently rely on delegate node specs not being regenerated,

# as the spec is set somewhat manually when adding the call delegate node.

# If we regenerate, it can change and break lowering (it becomes a tuple?).

# Ideally, we should figure out how to make the spec regeneration not break

# things.

#

# We do need to regenerate non-call-delegate node specs, as this pass is called

# multiple times in some lowering paths (backends can and do call it).

if "spec" not in node.meta:

node.meta["spec"] = pytree.tree_map(make_spec, meta_val)

else:

node.meta["spec"] = pytree.tree_map(make_spec, meta_val)

else:

if "spec" not in node.meta:

node.meta["spec"] = pytree.tree_map(make_spec, meta_val)

Sure - the issue that I've seen is that sometimes this pass gets called multiple times (Cadence backend does this, for example) and thus we need to regenerate the spec for most nodes to make sure they pick up on any shape changes between calls.

But if we regenerate the spec for call_delegate nodes, it breaks things. So the if "spec" not in node.meta: condition should only apply to call_delegate but not anything else. Otherwise is breaks existing backend assumptions.

Ideally, we'll do a deeper change to fix this but this preserves the existing behavior. I could change the line to if "spec" not in node.meta or node.target != executorch_call_delegate if you'd prefer.

Summary: Our current SpecPropPass doesn't properly capture the effect of guards in the shape environment due to double-tracing certain ops. The problem looks like this: * Every time we trace through the graph, we generate new symints. * That's fine, since shape_env will pick up guards during the retrace. * Problem is that SpecPropPass does this twice. Once to generate the spec and then once by calling super().call_operator(...) ([https://github.com/.../exir/passes/spec_prop_pass.py...](https://github.com/pytorch/executorch/blob/11f752cf84b296a39c0b74b889d618f279bc8186/exir/passes/spec_prop_pass.py#L98)). * The tensor spec gets the symint from the first. But the graph and guards use the second. * Hence the tensor spec doesn't pick up on guards. To resolve this, I've updated the SpecPropPass to re-trace the graph and then generate specs based on the meta values, not the traced ProxyValues (thanks angelayi for the suggestion). This resolves the issue. I originally saw this issue with the NMS torchvision op, but to avoid adding a new dep to the core EXIR tests, I've written a test with a custom op that uses an unbacked symint in the meta kernel output shape to replicate the bug in the same way. Differential Revision: D85913581 Pulled By: GregoryComer

GregoryComer requested review from JacobSzwejbka and larryliu0820 as code owners October 31, 2025 00:23

meta-cla bot added the CLA Signed This label is managed by the Facebook bot. Authors need to sign the CLA before a PR can be reviewed. label Oct 31, 2025

meta-codesync bot added fb-exported meta-exported labels Oct 31, 2025

GregoryComer marked this pull request as draft October 31, 2025 00:24

GregoryComer added ciflow/trunk release notes: exir Changes to any dialects and passes on these dialects, such as memory planning labels Oct 31, 2025

GregoryComer force-pushed the export-D85913581 branch from db9ef9c to 3944aac Compare November 6, 2025 23:42

pytorch-bot bot pushed a commit that referenced this pull request Nov 6, 2025

Fix shape_env handling in SpecPropPass (WIP) (#15485)

3944aac

Summary: Pull Request resolved: #15485 Differential Revision: D85913581

GregoryComer added a commit to GregoryComer/executorch that referenced this pull request Nov 6, 2025

Fix shape_env handling in SpecPropPass (WIP) (pytorch#15485)

6c9d1cb

Summary: Pull Request resolved: pytorch#15485 Differential Revision: D85913581

GregoryComer force-pushed the export-D85913581 branch from 3944aac to 6c9d1cb Compare November 6, 2025 23:49

GregoryComer added a commit to GregoryComer/executorch that referenced this pull request Nov 6, 2025

Fix shape_env handling in SpecPropPass (WIP) (pytorch#15485)

8213149

Summary: Pull Request resolved: pytorch#15485 Differential Revision: D85913581

GregoryComer force-pushed the export-D85913581 branch from 6c9d1cb to 8213149 Compare November 6, 2025 23:58

GregoryComer requested a review from angelayi November 7, 2025 00:09

GregoryComer changed the title ~~Fix shape_env handling in SpecPropPass~~ Fix double-tracing in SpecPropPass Nov 7, 2025

angelayi reviewed Nov 7, 2025

View reviewed changes

exir/passes/spec_prop_pass.py Outdated Show resolved Hide resolved

GregoryComer force-pushed the export-D85913581 branch from 8213149 to 01418d5 Compare November 8, 2025 02:44

GregoryComer added a commit to GregoryComer/executorch that referenced this pull request Nov 8, 2025

Fix shape_env handling in SpecPropPass (WIP) (pytorch#15485)

01418d5

Summary: Pull Request resolved: pytorch#15485 Differential Revision: D85913581

GregoryComer added a commit to GregoryComer/executorch that referenced this pull request Nov 10, 2025

Fix shape_env handling in SpecPropPass (WIP) (pytorch#15485)

0152b4a

Summary: Pull Request resolved: pytorch#15485 Differential Revision: D85913581

GregoryComer force-pushed the export-D85913581 branch from 01418d5 to 0152b4a Compare November 10, 2025 22:37

GregoryComer added a commit to GregoryComer/executorch that referenced this pull request Nov 11, 2025

Fix shape_env handling in SpecPropPass (WIP) (pytorch#15485)

83de667

Summary: Pull Request resolved: pytorch#15485 Differential Revision: D85913581

GregoryComer force-pushed the export-D85913581 branch from 0152b4a to 83de667 Compare November 11, 2025 00:42

GregoryComer marked this pull request as ready for review November 11, 2025 05:10

GregoryComer added a commit to GregoryComer/executorch that referenced this pull request Nov 12, 2025

Fix shape_env handling in SpecPropPass (WIP) (pytorch#15485)

19eef20

Summary: Pull Request resolved: pytorch#15485 Differential Revision: D85913581

GregoryComer force-pushed the export-D85913581 branch from 83de667 to 19eef20 Compare November 12, 2025 00:40

GregoryComer requested a review from angelayi November 12, 2025 01:32

GregoryComer added a commit to GregoryComer/executorch that referenced this pull request Nov 12, 2025

Fix shape_env handling in SpecPropPass (WIP) (pytorch#15485)

1996324

Summary: Pull Request resolved: pytorch#15485 Differential Revision: D85913581

GregoryComer force-pushed the export-D85913581 branch from 19eef20 to 1996324 Compare November 12, 2025 23:38

mergennachin requested a review from Copilot December 4, 2025 18:19

Copilot started reviewing on behalf of mergennachin December 4, 2025 18:20 View session

Copilot finished reviewing on behalf of mergennachin December 4, 2025 18:24

Copilot AI reviewed Dec 4, 2025

View reviewed changes

exir/passes/spec_prop_pass.py Show resolved Hide resolved

exir/passes/spec_prop_pass.py Show resolved Hide resolved

exir/tests/test_passes.py Outdated Show resolved Hide resolved

larryliu0820 reviewed Dec 4, 2025

View reviewed changes

GregoryComer force-pushed the export-D85913581 branch from 1996324 to 96e10d5 Compare December 4, 2025 20:52

GregoryComer force-pushed the export-D85913581 branch from 96e10d5 to dcf957f Compare December 4, 2025 22:07

GregoryComer force-pushed the export-D85913581 branch from dcf957f to b3cc905 Compare December 4, 2025 22:19

GregoryComer requested a review from larryliu0820 December 4, 2025 22:20

larryliu0820 reviewed Dec 4, 2025

View reviewed changes

larryliu0820 approved these changes Dec 4, 2025

View reviewed changes

GregoryComer force-pushed the export-D85913581 branch from b3cc905 to 7da6e25 Compare December 6, 2025 03:38

Fix double-tracing in SpecPropPass #15485

Are you sure you want to change the base?

Fix double-tracing in SpecPropPass #15485

Conversation

GregoryComer commented Oct 31, 2025 • edited Loading Uh oh! There was an error while loading. Please reload this page.

Uh oh!

Uh oh!

pytorch-bot bot commented Oct 31, 2025 • edited Loading Uh oh! There was an error while loading. Please reload this page.

Uh oh!

🔗 Helpful Links

🧪 See artifacts and rendered test results at hud.pytorch.org/pr/pytorch/executorch/15485

✅ You can merge normally! (1 Unrelated Failure)

Uh oh!

meta-codesync bot commented Oct 31, 2025

Uh oh!

GregoryComer commented Nov 7, 2025

Uh oh!

Uh oh!

meta-codesync bot commented Nov 11, 2025

Uh oh!

GregoryComer commented Nov 13, 2025

Uh oh!

Copilot AI left a comment

Choose a reason for hiding this comment

Pull request overview

Reviewed changes

Uh oh!

Uh oh!

Uh oh!

Uh oh!

larryliu0820 Dec 4, 2025

Choose a reason for hiding this comment

Uh oh!

GregoryComer Dec 4, 2025 • edited Loading Uh oh! There was an error while loading. Please reload this page.

Uh oh!

Choose a reason for hiding this comment

Uh oh!

larryliu0820 Dec 4, 2025

Choose a reason for hiding this comment

Uh oh!

GregoryComer Dec 4, 2025

Choose a reason for hiding this comment

Uh oh!

larryliu0820 Dec 4, 2025

Choose a reason for hiding this comment

Uh oh!

GregoryComer Dec 4, 2025

Choose a reason for hiding this comment

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

3 participants

GregoryComer commented Oct 31, 2025 •

edited

Loading

pytorch-bot bot commented Oct 31, 2025 •

edited

Loading

GregoryComer Dec 4, 2025 •

edited

Loading